NSF PAR Search | NSF Public Access Repository

Note: When clicking on a Digital Object Identifier (DOI) number, you will be taken to an external site maintained by the publisher. Some full text articles may not yet be available without a charge during the embargo (administrative interval).
What is a DOI Number?

Some links on this page may take you to non-federal websites. Their policies may differ from this site.

Contrastive Unlearning: A Contrastive Approach to Machine Unlearning

https://doi.org/10.24963/ijcai.2025/830

Lee, Hong kyu; Zhang, Qiuchen; Yang, Carl; Lou, Jian; Xiong, Li (September 2025, International Joint Conferences on Artificial Intelligence Organization)

Machine unlearning aims to eliminate the influence of a subset of training samples (i.e., unlearning samples) from a trained model. Effectively and efficiently removing the unlearning samples without negatively impacting the overall model performance is challenging. Existing works mainly exploit input and output space and classification loss, which can result in ineffective unlearning or performance loss. In addition, they utilize unlearning or remaining samples ineffectively, sacrificing either unlearning efficacy or efficiency. Our main insight is that the direct optimization on the representation space utilizing both unlearning and remaining samples can effectively remove influence of unlearning samples while maintaining representations learned from remaining samples. We propose a contrastive unlearning framework, leveraging the concept of representation learning for more effective unlearning. It removes the influence of unlearning samples by contrasting their embeddings against the remaining samples' embeddings so that their embeddings are closer to the embeddings of unseen samples. Experiments on a variety of datasets and models on both class unlearning and sample unlearning showed that contrastive unlearning achieves the best unlearning effects and efficiency with the lowest performance loss compared with the state-of-the-art algorithms. In addition, it is generalizable to different contrastive frameworks and other models such as vision-language models. Our main code is available on github.com/Emory-AIMS/Contrastive-Unlearning
more » « less
Free, publicly-accessible full text available September 1, 2026
DPAR: Decoupled Graph Neural Networks with Node-Level Differential Privacy

https://doi.org/10.1145/3589334.3645531

Zhang, Qiuchen; Lee, Hong kyu; Ma, Jing; Lou, Jian; Yang, Carl; Xiong, Li (May 2024, International World Wide Web Conference (WWW))

Full Text Available
Temporal Network Embedding via Tensor Factorization

https://doi.org/10.1145/3459637.3482200

Ma, Jing; Zhang, Qiuchen; Lou, Jian; Xiong, Li; Ho, Joyce C. (October 2021, Proceedings of the 30th ACM International Conference on Information & Knowledge Management)

Full Text Available
Communication Efficient Tensor Factorization for Decentralized Healthcare Networks

https://doi.org/10.1109/ICDM51629.2021.00147

Ma, Jing; Zhang, Qiuchen; Lou, Jian; Xiong, Li; Bhavani, Sivasubramanium; Ho, Joyce C. (December 2021, 2021 IEEE International Conference on Data Mining (ICDM))

Full Text Available
Spatio-Temporal Tensor Sketching via Adaptive Sampling

https://doi.org/10.1007/978-3-030-67658-2_28

Ma, Jing; Zhang, Qiuchen; Ho, Joyce C; Xiong, Li (February 2021, ECML PKDD 2020: Machine Learning and Knowledge Discovery in Databases)
null (Ed.)
Full Text Available
Communication Efficient Federated Generalized Tensor Factorization for Collaborative Health Data Analytics

https://doi.org/10.1145/3442381.3449832

Ma, Jing; Zhang, Qiuchen; Lou, Jian; Xiong, Li; Ho, Joyce C. (April 2021, Proceedings of the Web Conference 2021)
null (Ed.)
Full Text Available
Privacy-Preserving Tensor Factorization for Collaborative Health Data Analysis

https://doi.org/10.1145/3357384.3357878

Ma, Jing; Zhang, Qiuchen; Lou, Jian; Ho, Joyce C.; Xiong, Li; Jiang, Xiaoqian (November 2019, Proceedings of the 28th ACM International Conference on Information and Knowledge Management)

Tensor factorization has been demonstrated as an efficient approach for computational phenotyping, where massive electronic health records (EHRs) are converted to concise and meaningful clinical concepts. While distributing the tensor factorization tasks to local sites can avoid direct data sharing, it still requires the exchange of intermediary results which could reveal sensitive patient information. Therefore, the challenge is how to jointly decompose the tensor under rigorous and principled privacy constraints, while still support the model's interpretability. We propose DPFact, a privacy-preserving collaborative tensor factorization method for computational phenotyping using EHR. It embeds advanced privacy-preserving mechanisms with collaborative learning. Hospitals can keep their EHR database private but also collaboratively learn meaningful clinical concepts by sharing differentially private intermediary results. Moreover, DPFact solves the heterogeneous patient population using a structured sparsity term. In our framework, each hospital decomposes its local tensors and sends the updated intermediary results with output perturbation every several iterations to a semi-trusted server which generates the phenotypes. The evaluation on both real-world and synthetic datasets demonstrated that under strict privacy constraints, our method is more accurate and communication-efficient than state-of-the-art baseline methods.
more » « less
Full Text Available

Search for: All records